Papers with GazeVQA tasks
A Gaze-grounded Visual Question Answering Dataset for Clarifying Ambiguous Japanese Questions (2024.lrec-main)
Copied to clipboard
| Challenge: | Visual question answering (VQA) tasks are often based on directives, which can cause ambiguities in human utterances. |
| Approach: | They propose a method that clarifies ambiguous questions using gaze information . they propose combining gaze information with gaze information to improve accuracy . |
| Outcome: | The proposed method improves performance in some cases of a GazeVQA system on Gaze. |